The integration of information retrieval techniques within a software reuse environment

نویسندگان

  • Forbes Gibb
  • Colm McCartan
  • Ruairi O'Donnell
  • Niall Sweeney
  • Ruben Leon
چکیده

This paper describes the development of an information retrieval (IR) model for the indexing, storage and retrieval of documents created in extensible mark-up language (XML). The application area is the software reuse environment, which involves a broader class of documents than can be processed by conventional IR systems. This includes design and analysis documents in unified modelling language (UML) notation, as well as textual format, source code and textual and source code component interface definitions. XML was selected because it is emerging as the key standard for the representation of structured documents on the World Wide Web (WWW) and incorporates methods for the representation of metadata. A model is described that is easily customisable, since it is based upon an extensible object-oriented framework. This allows the development of an IR architecture that can easily be adapted to cope with the proliferation of XML document type definitions (DTDs) that is likely to be a characteristic of the WWW in the near future.

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Facilitating an Automated Approach to Architecture-based Software Reuse

Over the past several years, a number of techniques have been developed to address various issues involving software reuse, such as component classification, retrieval, and integration. However, it is not adequate to only have reuse techniques that address reuse issues separately. Instead, a seamless integration of these reuse techniques is critical to achieve effective reuse. In this paper, we...

متن کامل

Boosting Passage Retrieval through Reuse in Question Answering

Question Answering (QA) is an emerging important field in Information Retrieval. In a QA system the archive of previous questions asked from the system makes a collection full of useful factual nuggets. This paper makes an initial attempt to investigate the reuse of facts contained in the archive of previous questions to help and gain performance in answering future related factoid questions. I...

متن کامل

Investigating Embedded Question Reuse in Question Answering

The investigation presented in this paper is a novel method in question answering (QA) that enables a QA system to gain performance through reuse of information in the answer to one question to answer another related question. Our analysis shows that a pair of question in a general open domain QA can have embedding relation through their mentions of noun phrase expressions. We present methods f...

متن کامل

The use of Mediators for Component Retrieval in a Reuse Environment

CBD (Component Based Development) aims at constructing software through the inter-relationship between preexisting components. However, these components must be connected to a specific domain of application in order to reuse components effectively. In general, reusable components are stored in a great variety of data sources. Thus, a possible solution for accessing domain information is to use ...

متن کامل

An approach to improve the effect iveness of software retrieval

This paper introduces an approach for software retrieval based on partial natural language analysis of both queries and descriptions of software components. Retrieval is accomplished by matching the semantic structures associated to both the query and the natural language descriptions of software components in a knowledge base. Semantic structures are automatically created by identification of ...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

عنوان ژورنال:
  • J. Information Science

دوره 26  شماره 

صفحات  -

تاریخ انتشار 2000